HBase Reviews & Ratings 2024

Overview

What is HBase?

The Apache HBase project's goal is the hosting of very large tables -- billions of rows X millions of columns -- atop clusters of commodity hardware. Apache HBase is an open-source, distributed, versioned, non-relational database modeled after Google's Bigtable.

Recent Reviews

TrustRadius Insights

December 14, 2023

HBase has established itself as a crucial tool for various organizations, including PayPal, to store and retrieve records in near real …

HBase as the brother of big data

7 out of 10

March 16, 2019

Incentivized

I use HBase because it is a NoSQL database and it is open sourced and can store big data. We can store any structured, semi-structured and …

An Amazing Experience

9 out of 10

December 14, 2018

Incentivized

HBase was used as a datastore for the retails product catalog and session maangement

HBASE!!!

8 out of 10

December 13, 2018

Incentivized

HBase is used as part of the company's main revenue generating platform. We're using it store data with usages of mapreduce, generates …

No SQL Database to Support Near Real Time Analytics

8 out of 10

April 19, 2018

Incentivized

HBase was used in my previous organization(PayPal) where we needed a database for storing and retrieving records in near real time. It was …

HBase, The Only Enterprise NoSQL Choice

10 out of 10

April 06, 2018

Incentivized

HBase is being used by multiple organizations and internally it is used company-wide. it solves a large range of problems and provides …

HBase

10 out of 10

September 13, 2017

Incentivized

HBase solves problems of scalability and management of multi-terabyte applications. It makes scaling to +1 nodes very easy, especially …

HBase - good for UI performance from a Hadoop cluster

7 out of 10

April 14, 2017

Incentivized

HBase makes it possible to provide sub-second UI responsiveness when querying very large data sets. This is in contrast to something like H…

Support for HBase

7 out of 10

April 04, 2017

Incentivized

It is used as a data store that consolidates the updates from the upstream key-value store where upstream store only stores the updates …

HBase - a scaleable, consistent data store

7 out of 10

June 20, 2016

Incentivized

We use HBase as a secondary data store. We chose it mainly for its strongly consistent data model, and scalability. It has a pretty good …

Apache HBase: Through the Looking Glass!

7 out of 10

November 25, 2015

Incentivized

Apache HBase was used for mastering solutions, for creating master data sets and reconciling conflicting data coming to Apache Hadoop …

Read all reviews

Awards

Products that are considered exceptional by their customers based on a variety of criteria win TrustRadius awards. Learn more about the types of TrustRadius awards to make the best purchase decision. More about TrustRadius Awards

Popular Features

View all 7 features

Availability (5)
7.8
78%
Security (5)
7.8
78%
Performance (5)
7.1
71%
Concurrency (5)
7.0
70%

Return to navigation

Pricing

View all pricing

HBase

N/A

Unavailable

What is HBase?

Entry-level set up fee?

No setup fee

Offerings

Free Trial
Free/Freemium Version
Premium Consulting/Integration Services

Would you like us to let the vendor know that you want pricing?

23 people also want pricing

Alternatives Pricing

MarkLogic Server

$0.01

per MCU/per hour + 0.10 per GB/per month

What is MarkLogic Server?

MarkLogic Server is a multi-model database that has both NoSQL and trusted enterprise data management capabilities. The vendor states it is the most secure multi-model database, and it’s deployable in any environment. They state it is an ideal database to power a data hub.

Couchbase Capella

$0.32

per hour per node

What is Couchbase Capella?

The Couchbase Capella product is a fully managed DBaaS automating setup and ongoing operations.

Return to navigation

Product Demos

Apache Hbase Tutorial | Hadoop Hbase | Hbase Training | Intellipaat

YouTube

Return to navigation

Features

NoSQL Databases

NoSQL databases are designed to be used across large distrusted systems. They are notably much more scalable and much faster and handling very large data loads than traditional relational databases.

7.7

Avg 8.8

7.1
Performance
(5) Ratings
How fast the database performs under data load
7.8
Availability
(5) Ratings
Availability is the probability that the NoSQL database will be available to preform its function when called upon.
7
Concurrency
(5) Ratings
Concurrency is the ability for multiple processes to access or change shared data simultaneously. The greater the number of concurrent user processes that can execute without blocking each other, the greater the concurrency of the database system.
7.8
Security
(5) Ratings
Security features include authentication against external security mechanisms liker LDAP, Windows Active Directory, and authorization or privilege management. Some NoSQL databases also support encryption.
8.6
Scalability
(5) Ratings
NoSQL databases are inherently more stable than relational databases and have built-in support for replication and partitioning of data to support scalability.
7.1
Data model flexibility
(5) Ratings
NoSQL databases do not rely on rely on tables, columns, rows, or schemas to organize and retrieve data, but use use more flexible data models to accommodate the large volume and variety of data being generated by modern applications.
8.2
Deployment model flexibility
(5) Ratings
Can be deployed on-premise or in the cloud.

Return to navigation

Product Details

About
Tech Details
FAQs

What is HBase?

HBase Technical Details

Operating Systems	Unspecified
Mobile Application	No

Frequently Asked Questions

Reviewers rate Scalability highest, with a score of 8.6.

The most common users of HBase are from Mid-sized Companies (51-1,000 employees).

Return to navigation

Comparisons

View all alternatives

Compare with

Reviews and Ratings

(32)

January 31st 2024

Community Insights

TrustRadius Insights are summaries of user sentiment data from TrustRadius reviews and, when necessary, 3rd-party data sources. Have feedback on this content? Let us know!

Business Problems Solved
Recommendations

HBase has established itself as a crucial tool for various organizations, including PayPal, to store and retrieve records in near real time. Users have found that HBase excels in analytical use cases by providing faster lookup of records with consistent reads and writes, making it ideal for handling large datasets. It allows for faster querying of records compared to other NoSQL databases, resulting in improved data access and analysis capabilities. The ease of installation and configuration, thanks to its integration with the HDP Hortonworks stack, is another advantage that users appreciate.

One significant use case for HBase is as a data store for streaming data ingested through mechanisms like Apache NiFi, Apache Storm, Apache Spark Streaming, Apache Flink, and Streaming Analytics Manager. This allows organizations to efficiently manage and process continuous streams of data. Furthermore, HBase's ability to store structured, semi-structured, and unstructured data without requiring a pre-defined schema makes it a versatile choice for a range of applications.

Customers across industries have leveraged HBase successfully for their specific needs. In the retail sector, it serves as a datastore for product catalogs, session management systems, and revenue-generating platforms. Additionally, businesses involved in advertising and location analytics rely on HBase to generate locational information efficiently. Its scalability and read performance with avro data containing geospatial information make HBase preferable over alternatives like Cassandra.

HBase also plays a vital role in managing data within Apache Hadoop systems. It is used to create master data sets and reconcile conflicting data. Moreover, HBase serves as a secondary layer of storage that consolidates updates from upstream key-value stores.

While users highly recommend HBase for its data model consistency, scalability, and well-documented features, they do acknowledge the operational overhead associated with deploying and managing clusters. Nonetheless, this does not overshadow the significant benefits that organizations derive from using HBase to solve scalability and management issues related to multi-terabyte applications.

HBase is recommended for handling huge amounts of data and integrating with other tools. Users find HBase to be a good choice for scenarios requiring streaming ingest, fast lookups, and processing massive datasets. Its integration capabilities with other tools make it a valuable asset for organizations dealing with vast amounts of data.

HBase is also highly regarded for its real-time reporting capabilities over big data and seamless integration with business intelligence (BI) tools. Users recommend HBase as a reliable NoSQL datastore specifically designed to handle big data loads. It serves as an effective solution for storing unstructured or semi-structured data while providing easy integration with frontend applications.

Another common recommendation is to consider HBase's native integration with Hadoop and other data access engines. Users find HBase helpful for storing and processing non-relational data efficiently. Additionally, they recommend it as a reliable option for data storage and provision to other applications, making it suitable for various use cases.

It is important to evaluate HBase when considering NoSQL databases, as it offers unique benefits such as amazing structured/unstructured data storage capabilities and support for parallel programming. Users suggest utilizing HBase for specific use cases where large amounts of similar data need to be stored and accessed easily.

Lastly, users emphasize the importance of proper data modeling and workload tuning for successful implementation of HBase. They advise against using HBase for full table scan workloads and suggest considering relational databases when applicable. Additionally, they encourage the use of HBase for OLAP and OLTP use cases, highlighting its suitability for handling huge datasets and analytical processing needs.

Attribute Ratings

7.9
Likelihood to Renew10 ratings

Reviews

(1-4 of 4)

Sort By *

Companies can't remove reviews or game the system. Here's why

December 13, 2018

HBASE!!!

Anson Abraham

Data Lord

Envisagenics, Inc. (Marketing and Advertising, 51-200 employees)

Score 8 out of 10

Vetted Review

Verified User

Incentivized

Use Cases and Deployment Scope

HBase is used as part of the company's main revenue generating platform. We're using it store data with usages of mapreduce, generates locational information for advertising business and location analytics. Storage wise, it made sense to use HBASE over Cassandra, as well as for read performance with avro data with geospatial information in the data

Pros and Cons

Excellent for read performance
Great store of file format of avro
Easy integration into mapreduce
Replication ability

Write performance
Performance support for parquet file format. supports, but performance wise still not there
API / library availability for spark, rather than creating a new library for it

HBase Feature Ratings

NoSQL Databases (7)

60%

6.0

Performance: 50%
5.0
Availability: 50%
5.0
Concurrency: 30%
3.0
Security: 60%
6.0
Scalability: 70%
7.0
Data model flexibility: 80%
8.0
Deployment model flexibility: 80%
8.0

Likelihood to Recommend

It does depend on the use case scenario. It works really well if your schema doesn't really need relational features. It's really good for that. If you want to run as transactional, not a good idea. Relational analytics is not good for this, as well as edge network data. If you're using PB of data, then HBASE is best suited in this case as well.

Return on Investment

Negative ROI has been on hardware usage. When used frequently, we have had constant disk failures. As a result, it requires HDD replacements.
But with disk failures, HA is available, however, to a certain extent.
Large datasets helped causality issues to be mitigated.

Alternatives Considered

Cassandra

Cassandra os great for writes. But with large datasets, depending, not as great as HBASE. Cassandra does support parquet now. HBase still performance issues. Cassandra has use cases of being used as time series. HBase, it fails miserably. GeoSpatial data, Hbase does work to an extent. HA between the two are almost the same.

Other Software Used

Cassandra, Apache Solr, Elasticsearch, Apache Spark, PostgreSQL, MariaDB, Amazon DynamoDB, Azure Cosmos DB

Likelihood to Renew

Hbase is open source. So will be using it in any case. If it was made into commercial product, strong possibility of not using HBase, and would probably use something else at that point, most likely Cassandra. HBase does scale, if done correctly, and will perform if used correctly. Would reocmmend to use.

April 19, 2018

No SQL Database to Support Near Real Time Analytics

Vinaybabu Raghunandha Naidu

Software Engineer - Big Data Platform

Wish - Shopping Made Fun! (Computer Software, 201-500 employees)

Score 8 out of 10

Vetted Review

Verified User

Incentivized

Use Cases and Deployment Scope

HBase was used in my previous organization(PayPal) where we needed a database for storing and retrieving records in near real time. It was used within consumer analytics and other sub-teams. It supported our near real-time use analytical cases by proving a faster lookup of records with consistency reads/writes. Apart from that, helped in querying the records much faster than other NoSQL databases.

Pros and Cons

Faster lookup of records using the row keys. It helped to fetch thousands of records in a much faster way using the row keys
As it is a columnar data store, helped us to improve the query performance and aggregations
Sharding helps us to optimize the data storage and retrieval. HBase provides automatic or manually sharding of tables.
Dynamic addition of columns and column family helped us to modify the schema with ease.

Identified issues with Hmaster when handling a huge number of nodes
Cannot have multiple indexes as row key is the only column which could be indexed.
HBase does not support partial row keys which limit its query performance.

HBase Feature Ratings

NoSQL Databases (7)

75.71428571428571%

7.6

Performance: 80%
8.0
Availability: 70%
7.0
Concurrency: 80%
8.0
Security: 60%
6.0
Scalability: 60%
6.0
Data model flexibility: 90%
9.0
Deployment model flexibility: 90%
9.0

Likelihood to Recommend

Hbase is well suited for large organizations with millions of operations performing on tables, real-time lookup of records in a table, range queries, random reads and writes and online analytics operations.

Hbase cannot be replaced for traditional databases as it cannot support all the features, CPU and memory intensive. Observed increased latency when using with MapReduce job joins.

Return on Investment

It supports the near real-time use cases when integrated with Spark Streaming.
It helps to store huge volume of records with consistent reads/writes.
Maintenance is the pain point as it requires some maintenance and monitoring of regional servers and nodes

Alternatives Considered

MongoDB, MySQL and Teradata Database

Compared NoSQL databases with traditional databases for faster retrieval and consistency. As MongoDB is a NoSQL supports dynamic fields, however, query performance is bad for aggregations and added maintenance. When compared with MySQL and Teradata, it could not scale up as fast as Hbase and added cost involved to it. HBase can be easily scalable to a huge volume of records, have a faster lookup and provides consistency

Other Software Used

Apache Hive, Apache Spark, Looker

Likelihood to Renew

Compared to other MySQL databases Hbase provides better query results and dynamic in nature. It can be integrated to different computational engines like spark to support to real time use cases.

September 13, 2017

HBase

Verified User

Engineer in Information Technology

Computer Software Company, 501-1000 employees

Score 10 out of 10

Vetted Review

Reseller Incentivized

Use Cases and Deployment Scope

HBase solves problems of scalability and management of multi-terabyte applications. It makes scaling to +1 nodes very easy, especially through Ambari. It is built with fault tolerance and availability in mind. You can use it on a single node but it shines on multi-node infrastructure. With high data access speed and resiliency, I wouldn't recommend any other NoSQL database for general use.

Pros and Cons

HBase data access and retrieval only gets better with larger scale.
Fault tolerance is built in, if you have unreliable hardware, HBase will make every effort to keep your data online.
Extremely fast key lookups and write throughput.

Multi-tenancy is still work in progress
Usability and beginner friendliness
It has a bad reputation of being complex

Likelihood to Recommend

HBase typically fits well in low latency, tight SLA scenarios. It is not recommended to be used in situations where a relational database would fit better. So in essence, if you're trying to do a lot of analytical workloads or joins, HBase wouldn't fit so well. If primary key access is sufficient, then HBase is a good fit.

Return on Investment

We were able to scale our application from 5TB running in a relational database to 20TB on top of HBase
Application availability was always high even with half of the nodes having hardware issues
HBase can be used with standard mechanical harddrive storage. There's no need for fancy SAN or NAS storage with HBase which is almost always expensive.

Alternatives Considered

Cassandra and MongoDB

Typically, Cassandra is faster on reads and HBase is faster on writes. You use Cassandra when you want to use a website, HBase is just an overall good general use database engine. Cassandra has its own storage engine and HBase uses HDFS and all its benefits. MongoDB is typically also used in web development, it has a great support for JSON but it's been known for poor scalability. It also uses its own storage engine.

Likelihood to Renew

There's really not anything else out there that I've seen comparable for my use cases. HBase has never proven me wrong. Some companies align their whole business on HBase and are moving all of their infrastructure from other database engines to HBase. It's also open source and has a very collaborative community.

November 25, 2015

Apache HBase: Through the Looking Glass!

Rekha Joshi

Staff Software Engineer

Intuit (Computer Software, 5001-10,000 employees)

Score 7 out of 10

Vetted Review

Verified User

Incentivized

Use Cases and Deployment Scope

Apache HBase was used for mastering solutions, for creating master data sets and reconciling conflicting data coming to Apache Hadoop systems.

Pros and Cons

Apache HBase is a widely used java based distributed NoSQL environment on Apache Hadoop.
While there has been growing interest and efforts in in memory computing, there are investments on Apache Hadoop (or hadoop provider variants) across domains. So that is a large market.
I worked on HBase for applications which needed to provide strong consistency and interact with Apache Hadoop.

You could encounter issues like region is not online or NotServingException or region server going down, out of memory errors.
As HBase works with Zookeeper, care needs to be taken it is correctly set up. Most issues pertain usually to environment setup, configuration, shared load on system or maintenance.
The performance across workloads when evaluated against other NoSQL variants was not best in class, this is most times okay, but can be improved.
If you use Apache HBase, and want to upgrade it for some features then you might need to do a compatibility check against your Apache Hadoop and Apache HBase versions, there are dependency to think about.
The HBase master slave becomes the single point of failure, and may not be a preferred design.It is not highly available system.
Last I checked it did not have well tested easy integrations with Spark, and that can help.

Likelihood to Recommend

The key questions I ask when choosing NoSQL distributed database:

What is the application's inherent need? Does this component fit well in the design?
Does it provide high data security?
How does it assure there is no data loss?
How can I make sure it is a highly available system, and no downtime for customer?
Does it give me the best linear scalability?
What kind of tuning parameters does it allow the user to configure?
How does it stack up against other NoSQL variants on features, scalability, ease of use/contribute to and maturity of product?
What throughput can it attain under different kinds of workloads?

Return on Investment

Faster data insights
Better customer service

Return to navigation

HBase

MarkLogic Server

Couchbase Capella

Apache Hbase Tutorial | Hadoop Hbase | Hbase Training | Intellipaat

Elasticsearch

MongoDB

Redis™*

Couchbase Server

Amazon DynamoDB

IBM Cloudant

Cassandra

Azure Cosmos DB

Google Cloud Datastore

CouchDB

Community Insights